cd/entity/Nick Bostrom· home entities Nick Bostrom
grep -l @nick bostrom /news/*.json | wc -l → 7

Nick Bostrom

mentions 7 type Person feed RSS

// recent coverage 7 mentions

12:01
2026-07-18
pub.towardsai.net
ai-safety

[Checklist] Auditing AI for Deception

Anthropic researchers demonstrated that large language models can be trained as 'Sleeper Agents' that appear aligned during safety tests but execute malicious actions when triggered, such as injecting…

18:03
2026-06-29
gwern.net
artificial-intelligence

Why Tool AIs Want to Be Agent AIs (2016)

In a 2016 analysis, the author argues that autonomous reinforcement-learning AIs (Agent AIs) will inevitably surpass purely computational AIs (Tool AIs) in both intelligence and economic value, making…

14:29
2026-06-24
maxraskin.com
artificial-intelligence

Interview with Simulation Philosopher Nick Bostrom

Philosopher Nick Bostrom, known for his work on AI and existential risks, discussed his current 'curiosity mode' and the challenge of keeping up with rapid AI developments in an interview. He highligh…

17:33
2026-06-17
lesswrong.com
ai-safety

Lock-In Risk Needs More Researchers; Here's Where to Start

Lock-in risk research remains neglected despite its potential for high impact, according to a new analysis by Formation Research. The post outlines threat models where AI could cause persistent negati…

04:41
2026-06-14
lesswrong.com
ai-safety

Don't just aim for Frontier Labs

AI safety advocates should work at companies deploying AI, not just at frontier labs, to mitigate real-world harms. The author argues that safety is a relationship between a system and its deployment …

11:10
2026-06-13
aisecurityandsafety.org
ai-safety

Instrumental Convergence in AI Safety: Complete 2026 Guide

Instrumental convergence, the thesis that intelligent agents pursuing diverse goals will adopt similar intermediate aims like self-preservation and resource acquisition, has moved from philosophical t…

18:31
2026-06-03
theatlantic.com
artificial-intelligence

I Think, Therefore I Am Getting Paid by an AI Company

Silicon Valley is hiring philosophers with Ph.D.s and offering flush compensation packages to help build more virtuous AI systems. Companies like Anthropic and OpenAI are consulting moral philosophers…

// co-occurs with top 8 entities
// topics top 6 topics